Papers with speech recognition technology

2 papers
Evaluation Phonemic Transcription of Low-Resource Tonal Languages for Language Documentation (L18-1)

Copied to clipboard

Challenge: Language documentation involves recording the speech of native speakers.
Approach: They propose to use a neural network architecture to model phonemes and tones versus modelling them separately.
Outcome: The proposed method improves efficiency, minimizes typographical errors and maintains transcription faithfulness to acoustic signal while highlighting phonetic and phonemic facts for linguistic consideration.
Adversarial Speech Generation and Natural Speech Recovery for Speech Content Protection (2022.lrec-1)

Copied to clipboard

Challenge: Currently, researchers focus on how to protect the speaker's identifiable information, represented as voiceprint, contained in the speech.
Approach: They propose a frame-by-frame adversarial speech generation system to protect speech . they build an adversarials-based method that converts adversarially generated speech to human speech.
Outcome: The proposed method can encode and recover any sensitive audio, and it is easy to be conducted with publicly available speech recognition technology.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations